Papers with Wikipedia abstracts

2 papers
Bootstrapping Generators from Noisy Data (N18-1)

Copied to clipboard

Challenge: Existing methods for data-to-text generation focus on learning correspondences between structured data and associated texts.
Approach: They aim to bootstrap generators from large scale datasets where data and related texts are loosely aligned.
Outcome: The proposed model improves on a vanilla encoder-decoder which relies on soft attention.
T-REx: A Large Scale Alignment of Natural Language with Knowledge Base Triples (L18-1)

Copied to clipboard

Challenge: Existing datasets that provide alignments between natural language and knowledge bases (KB) triples are limited in size, lack coverage and are of unreported quality.
Approach: They propose to build a large scale dataset of alignments between Wikipedia abstracts and Wikidata triples that is two orders of magnitude larger than the largest available alignments dataset.
Outcome: The proposed dataset is two orders of magnitude larger than the largest available dataset and covers 2.5 times more predicates.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations